NSF PAR Search | NSF Public Access Repository

Note: When clicking on a Digital Object Identifier (DOI) number, you will be taken to an external site maintained by the publisher. Some full text articles may not yet be available without a charge during the embargo (administrative interval).
What is a DOI Number?

Some links on this page may take you to non-federal websites. Their policies may differ from this site.

Descent with Misaligned Gradients and Applications to Hidden Convexity

Bhaskara, Aditya; Cutkosky, Ashok; Kumar, Ravi; Purohit, Manish (April 2025, Proceedings of the 13th International Conference on Learning Representations (ICLR))

Full Text Available
Descent with Misaligned Gradients and Applications to Hidden Convexity

Bhaskara, Aditya; Cutkosky, Ashok; Kumar, Ravi; Purohit, Manish (April 2025, Proceedings of the 13th International Conference on Learning Representations (ICLR))

Full Text Available
Fully Unconstrained Online Learning

Cutkosky, Ashok; Mhammedi, Zakaria (December 2024, Advances in Neural Information Processing Systems)

Full Text Available
Adam with model exponential moving average is effective for nonconvex optimization

Ahn, Kwangjun; Cutkosky, Ashok (December 2024, Advances in Neural Information Processing Systems)

Full Text Available
The Road Less Scheduled

Defazio, Aaron; Xingyu, Alice Yang; Khaled, Ahmed; Mishchenko, Konstantin; Mehta, Harsh; Cutkosky, Ashok (December 2024, Advances in Neural Information Processing Systems)

Full Text Available
Online Linear Regression in Dynamic Environments via Discounting

Jacobsen, Andrew; Cutkosky, Ashok (July 2024, International Conference on Machine Learning)

Full Text Available
Random Scaling and Momentum for Non-smooth Non-convex Optimization

Zhang, Qinzi; Cutkosky, Ashok (July 2024, International Conference on Machine Learning)

Full Text Available
Private Zeroth-Order Nonsmooth Nonconvex Optimization

Zhang, Qinzi; Tran, Hoang; Cutkosky, Ashok (May 2024, International Conference on Learning Representations)

Full Text Available
Mechanic: A Learning Rate Tuner

Cutkosky, Ashok; Defazio, Aaron; Mehta, Harsh (December 2023, Advances in neural information processing systems (NeurIPS))

We introduce a technique for tuning the learning rate scale factor of any base optimization algorithm and schedule automatically, which we call Mechanic. Our method provides a practical realization of recent theoretical reductions for accomplishing a similar goal in online convex optimization. We rigorously evaluate Mechanic on a range of large scale deep learning tasks with varying batch sizes, schedules, and base optimization algorithms. These experiments demonstrate that depending on the problem, Mechanic either comes very close to, matches or even improves upon manual tuning of learning rates.
more » « less
Full Text Available
Unconstrained Online Learning with Unbounded Losses

Jacobsen, Andrew; Cutkosky, Ashok (July 2023, Proceedings of Machine Learning Research)

« Prev Next »

Search for: All records